Papers with classification algorithms
Augmenting Small Data to Classify Contextualized Dialogue Acts for Exploratory Visualization (2020.lrec-1)
Copied to clipboard
| Challenge: | a new corpus of conversations is being developed to support data visualization exploration . we use data augmentation to improve our methods for dialogue act classification . |
| Approach: | They propose to use a corpus of conversations to annotate contextualized dialogue acts . they highlight how thinking aloud affects interpretation of dialogue acts in the context . |
| Outcome: | The proposed AI can support visualization exploration with a small corpus of conversations . the proposed AI outperforms existing models in terms of performance and performance . |
An Annotated Dataset of Discourse Modes in Hindi Stories (2020.lrec-1)
Copied to clipboard
Swapnil Dhanwal, Hritwik Dutta, Hitesh Nankani, Nilay Shrivastava, Yaman Kumar, Junyi Jessy Li, Debanjan Mahata, Rakesh Gosangi, Haimin Zhang, Rajiv Ratn Shah, Amanda Stent
| Challenge: | Using a new corpus of sentences from Hindi short stories, we analyze the annotations for five different discourse modes argumentative, narrative, descriptive, dialogic and informative. |
| Approach: | They propose to annotate sentences from Hindi short stories for five different discourse modes argumentative, narrative, descriptive, dialogic and informative. |
| Outcome: | The proposed corpus has a high inter-annotator agreement (0.87 k-alpha) and is able to capture the nuances of the embedded discourse structures. |
Extracting Linguistic Knowledge from Speech: A Study of Stop Realization in 5 Romance Languages (2022.lrec-1)
Copied to clipboard
| Challenge: | voicing alternation phenomena of stops are a common problem in connected speech . phoneticians and phonologists are interested in analyzing phonetic variation . |
| Approach: | They use forced alignment with pronunciation variants and machine learning techniques to examine voicing alternations of stops in Romance languages. |
| Outcome: | The proposed method enables linguists to use large corpora and speech recognition systems . the results show that voicing alternations occur in all Romance languages . |
Crowdsourcing Regional Variation Data and Automatic Geolocalisation of Speakers of European French (L18-1)
Copied to clipboard
Jean-Philippe Goldman, Yves Scherrer, Julie Glikman, Mathieu Avanzi, Christophe Benzitoun, Philippe Boula de Mareüil
| Challenge: | a crowdsourcing platform is used to collect linguistic data and document language use, with a focus on regional variation in European French. |
| Approach: | They propose a crowdsourcing platform to collect linguistic data and document language use with a special focus on regional variation in European French. |
| Outcome: | The proposed platform collects linguistic data and documents language use with a special focus on regional variation in European French. |